Back

Pigment Cell & Melanoma Research

Wiley

Preprints posted in the last 7 days, ranked by how well they match Pigment Cell & Melanoma Research's content profile, based on 11 papers previously published here. The average preprint has a 0.01% match score for this journal, so anything above that is already an above-average fit.

1
Tripelennamine 1% topical cream versus diphenhydramine 1% topical cream for the relief of histamine-induced itching: A proof-of-concept study in healthy adults.

Nornoo, A. O.; Maarsingh, H.

2026-09-04 dermatology 10.64898/2026.09.01.26361939 medRxiv
Top 0.2%
1.0%
Show abstract

Introduction: There is an unmet need for effective topical anti-pruritic medications for acute itch, as there are only a few over-the-counter products that have a direct effect on itch. Tripelennamine is a first-generation antihistamine that would be useful in treating histamine-induced pruritus, however, supportive robust clinical data is lacking. Objectives: The efficacy of tripelennamine (TPA) compared to diphenhydramine (DPH) and a vehicle control cream base on histamine-induced pruritus was evaluated as the primary endpoint. Histamine-induced urticaria served as the secondary endpoint. Methods: Thirty-six healthy participants completed this single-center, double-blinded, placebo-controlled crossover clinical study. Following pretreatment with TPA1%, DPH 1% or vehicle control creams, histamine challenge occurred via iontophoresis and a visual analog scale (VAS) for pruritus was used to determine extent of itch (AUC-VAS), peak itch, and duration of itch. Results: Compared to the vehicle control, TPA reduced histamine-induced extent of itch (AUC-VAS), peak itch, and itch duration by 59%, 38% and 43%, respectively (p<0.01 all). DPH did not significantly affect these responses and TPA was superior in reducing extent of itch (48% reduction, p<0.05) and duration (38% shorter, p<0.05). TPA, but not DPH, also reduced histamine induced flare and wheal responses (secondary endpoints) by 53% and 27%, respectively. The reduction in flare responses by TPA was superior to that of DPH (45% reduction, p<0.05). Conclusion: TPA significantly attenuated histamine-induced pruritus and urticaria in a human histamine-challenge model and demonstrated greater efficacy than DPH. These findings provide strong evidence of the antipruritic activity of topical TPA and support further clinical investigation of TPA as a treatment for histaminergic itch and related dermatologic conditions.

2
Skin Cancer Classification Using Explainable Artificial Intelligence With an Ensemble Model and Rigorous Leakage Free Validation

BARAN, M. T.; KARAKOYUN, O.

2026-09-05 health systems and quality improvement 10.64898/2026.09.02.26362011 medRxiv
Top 0.5%
0.3%
Show abstract

Background: Reliable melanoma classification requires models that capture both local dermoscopic morphology and broader contextual patterns while maintaining auditable, leakageaware internal validation. Objectives: To develop and internally validate an EfficientNetB0-Swin Transformer Tiny ensemble for classifying histopathologically verified dermoscopic images as benign melanocytic lesions or malignant melanoma. Methods: This retrospective diagnostic model-development and internal validation study screened 552,869 ISIC Archive records; filtering and dermatologist review yielded 1,199 uniquepatient and unique lesion images (578 benign and 621 malignant). Images were the predictors and histopathology was the reference. ImageNet pretrained EfficientNetB0 and Swin-T features were fused. Patient independent five fold validation used weighted sampling, mixup, label smoothing, AdamW, early stopping, and five view test time augmentation. Results: Mean accuracy was 0.89325 {+/-} 0.03179, mean receiver operating characteristic area under the curve (ROC-AUC) was 0.96348 {+/-} 0.01695, and mean support weighted F1-score was 0.89300 {+/-} 0.03220. The fold level 95% confidence intervals were 0.8538-0.9327 for accuracy and 0.9424-0.9845 for ROC-AUC. Pooled counts were 526 true negatives, 52 false positives, 76 false negatives, and 545 true positives, yielding 87.76% sensitivity and 91.00% specificity. Qualitative Grad-CAM review showed peripheral artifact activation in two false positives and lesion centered activation in two correctly classified cases; these observations were not systematically scored. Limitations: The validation folds were also used for early stopping and checkpoint selection. Device stratified analysis, systematic interpretability scoring, calibration, and independent external validation were unavailable. Conclusions: The ensemble showed high internal discrimination and is intended only as a clinician facing adjunct. The error audit workflow enables targeted retrospective review, but external validation is required before clinical use or generalizability claims.

3
Spatial Geometry and Prevalence of Tunneling and Undermining in Pressure Ulcers

Frade, S.; Tunyiswa, Z.; Shin, M.; Dirks, R.

2026-09-01 dermatology 10.64898/2026.08.28.26361615 medRxiv
Top 0.6%
0.2%
Show abstract

Background: Pressure ulcers often develop complex three-dimensional morphologies that extend beyond the visible wound surface. Subsurface extensions such as tunneling and undermining create hidden cavities that complicate clinical assessment and wound management. Despite their clinical relevance, the prevalence and spatial characteristics of these subsurface wound morphologies have not been well characterized at scale. Methods: We performed a registry-based analysis using data from the LIFT-OFF Pressure Ulcer Registry, which captures longitudinal clinical documentation of pressure ulcers treated in routine care. The registry included approximately 18,000 patients with 32,000 documented pressure ulcers. Spatial characteristics of tunneling and undermining were analyzed using measurements recorded during routine wound assessments, including tract length, direction, and circumferential extent. Directional and circumferential distributions of subsurface defects were examined to characterize wound geometry. Results: Tunneling was present in 764 of 14,700 full-thickness pressure ulcers (5.2%), whereas undermining occurred in 2,293 wounds (15.6%). Tunneling tracts were typically short and exhibited directional clustering relative to the wound bed. In contrast, undermining demonstrated broader circumferential distributions and frequently involved larger subsurface separations beneath the wound margin. Both morphologies demonstrated distinct spatial patterns across anatomical locations and wound stages. Conclusion: Tunneling and undermining are common subsurface features of pressure ulcers and exhibit distinct spatial geometries. Whereas tunneling manifests as directional tract-like extensions, undermining more frequently produces circumferential tissue separation beneath wound margins. Improved characterization of subsurface wound architecture may enhance assessment of wound complexity and provide information not captured by surface measurements alone. Future studies should evaluate whether these features contribute to wound severity assessment, prognosis, and risk stratification.

4
A Randomized Non-Inferiority Trial of an eHealth Delivery Alternative for Cancer Genetic Testing for Hereditary Cancer (eREACH2)

Lee, K. T.; Egleston, B.; Fetzer, D.; Domchek, S. M.; Fleisher, L.; Wen, K.-Y.; Wagner, L.; Roberts, S.; Howe, S.; Cacioppo, C.; Christiansen, J.; Karpink, K.; Selmani, E.; Mastaglio, E.; Weinberg, M.; Wood, E. M.; Feng, J.; John, S.; Schweickert, K.; Mcleod, B.; Bradbury, A. R.

2026-09-03 genetic and genomic medicine 10.64898/2026.09.01.26361920 medRxiv
Top 0.9%
0.1%
Show abstract

Background: Many at-risk patients lack access to genetic services due to a genetic counselor (GC) workforce shortage. Little is known about how digital alternatives impact patients with and without cancer who meet criteria for genetic testing. Methods: eREACH2 is a randomized 4-arm non-inferiority trial where pre-test (visit 1) and/or return of results (visit 2) GC counseling was replaced with a patient-centered digital intervention. Arms include: A (GC/GC), B (GC/digital), C (digital/GC) and D (digital/digital). Primary outcomes were non-inferiority in uptake of genetic services and change in genetic knowledge and general anxiety from baseline to post-disclosure of results (T0-T2). Secondary cognitive and affective outcomes were assessed using non-inferiority ANOVAs and equivalency chi-squared tests in intention-to-treat and per-protocol analyses. Findings: 773 participants were recruited nationwide; 46.6% from rural areas. Mean age was 51 years (range 20-87), 13% male, 12% non-white, 29% had less than a college education, and 33% had a personal history of cancer. 584 (76%) patients completed testing (14% had a positive result, 16% had a VUS). In the primary ITT analyses, we met the non-inferiority for uptake of genetic services and anxiety, but results were inconclusive for knowledge. Secondary outcomes were heterogeneous across arms. Arm C demonstrated consistently favorable effects, while Arms B and D showed less favorable outcomes in select domains (e.g. satisfaction and MICRA). Patients who received positive or VUS results via digital disclosure had significantly higher MICRA scores - indicating greater negative response to testing. Interpretation: In this large, randomized trial of patients with and without cancer, the eREACH intervention was effective for pre-test counseling, but inconclusive for digital disclosure of results. Exploratory analyses suggest that digital delivery could be a reasonable alternative for individuals receiving negative results, while those receiving positive or VUS results may derive some short-term psychosocial benefit from GC disclosure.

5
Sedation Differentially Affects Distortion-Product And Stimulus-Frequency Otoacoustic Emissions In Chinchillas

Hauser, S. N.; Sivaprakasam, A. N.; Bharadwaj, H.; Heinz, M. G.

2026-09-01 physiology 10.64898/2026.08.26.746474 medRxiv
Top 0.9%
0.1%
Show abstract

Purpose: Otoacoustic emissions (OAEs) are used to assess outer hair cell (OHC) function. Clinical interpretation of OAE responses, however, is often limited to a present/absent binary since both physiological factors and measurement variability affect the measured OAE amplitude. Prior work showed elevated OAE responses in sedated compared to awake chinchillas, pointing to the potential influence of the medial olivocochlear (MOC) efferents on amplitudes, but this finding is inconsistent across species and OAE type. Here, we aimed to further investigate the effect of anesthesia on distortion- and reflection-type emissions in chinchillas using swept stimuli and more reliable calibration methods. Methods: Swept distortion-product (DP) and stimulus-frequency (SF) OAEs were measured in chinchillas with and without ketamine/xylazine sedation. Stimuli were presented using in-ear forward pressure level calibrations. DPOAE and SFOAE amplitudes and estimated Qerb from SFOAE group delays were compared across the two conditions. Results: We found that low-frequency DPOAE amplitudes were elevated when animals were sedated. The difference in SFOAE amplitudes was more variable across animals but appeared mildly reduced in sedated animals. Qerb estimates were slightly higher in sedated animals at some frequencies. The effect of sedation was not different across sexes. Conclusion: Taken together, these findings suggest that sedation impacts OAE measurements in chinchillas. MOC modulation could account for the present findings and differences across species. For diagnostic precision, OAE responses should be considered in the context of not only intrinsic OHC function but also extrinsic physiological processes that can modulate OHCs.

6
Conditional Myeloid-Specific Inhibition of UBE2N Hinders YUMM1.7 Growth

Schiavone, K.; Pecoraro, A.; Khawar, A.; Zhang, K.; Starczynowski, D.; Zhang, J. Y.

2026-09-01 cancer biology 10.64898/2026.08.31.748234 medRxiv
Top 1.0%
0.1%
Show abstract

The role of UBE2N in myeloid cell-mediated immune suppression in cancer remains undefined. Here, we examined the function of UBE2N in myeloid cell-mediated tumor progression using a temporally inducible myeloid-specific knockout model (LysMCreERUbe2nfl/fl). Temporally induced deletion of Ube2n in myeloid cells (Ube2nMyeKO) significantly hindered growth of YUMM1.7 melanoma. This was accompanied by reduced myeloid cell burden within the tumor microenvironment. We observed altered abundance of PD-1, PD-L1, and SPP1 in the Ube2nMyeKO tumor microenvironment at the tissue level. In vitro analysis showed that knock-in expression of a catalytically deficient UBE2NC87S mutant in bone marrow-derived macrophages (BMDMs) markedly decreased expression of Spp1. We observed decreased SPP1 secretion in Ube2nMyeKO BMDM-conditioned media (CM). Treatment with Ube2nMyeKO BMDM-CM decreased co-expression of PD-1, TIM-3, and LAG-3 on chronically stimulated T cells. Antibody-mediated neutralization of SPP1 in Ube2nWT BMDM-CM decreased PD-1 expression on CD8+ T cells. Together, these findings suggest a role for myeloid UBE2N in YUMM1.7 progression.

7
A germline KDM3C polymorphism impairs DNA repair and sensitizes to chemoradiotherapy

Hasan, A.; Demidova, E. V.; Priyadarshini, P.; Czyzewicz, P.; Gathuka, L.; Murayama, T.; Zhou, Y.; Kiss, Z. A.; Shastry, R. K.; Andrake, M.; Hearne, G.; Devarajan, K.; Wu, C.; Shah, A.; Schultz, B. M.; Connolly, D. C.; Rosen, G. L.; Canadas, I.; Liu, J. C.; Burtness, B. A.; Smith, J. J.; Dunbrack, R. L.; Golemis, E. A.; Whetstine, J. R.; Meyer, J. E.; Arora, S.

2026-08-31 genetic and genomic medicine 10.64898/2026.08.26.26360896 medRxiv
Top 1%
0.1%
Show abstract

Chemoradiotherapy (CRT) is the standard-of-care therapy for many solid malignancies, yet predictive biomarkers of treatment response remain limited. We identified a germline single nucleotide polymorphism (SNP) in an intrinsically disordered region of the lysine demethylase KDM3C/JMJD1C (p.S464T) that is associated with CRT outcomes in locally advanced rectal cancers (LARC) and head and neck squamous cell carcinoma (LA-HNSCC). In silico modeling with AlphaFold predicted S464T substitution influenced interaction between phosphorylated KDM3C and RNF8 FHA domain. In cellular models, conversion of S464 to T464 increased sensitivity to DNA-damaging agents. S464T substitution impaired damage-induced MDC1-RAP80 signaling and downstream RAP80-BRCA1 colocalization. SNP carrying cells impaired DNA repair causing genotoxic stress that is associated with increased cGAS-cGAMP innate immune signaling and increased apoptosis. Population analyses with the SNP highlighted an increase incidence of UV-induced skin and other cancers, linking inherited variation in the chromatin regulatory gene KDM3C to genome instability, cancer risk, and therapeutic vulnerability.

8
GLP-1 Receptor Agonist Initiation and Anti-VEGF Treatment Frequency in Diabetic Macular Edema: an IRIS(R) Registry Cohort Study

Nagalamadaka, P.; Ross, C. J.; Gilbert, J. B.; Stillman, H.; Ghauri, S. Y.; Dutton, S. M.; Kearney, W.; Li, J. H.; Leong, A.; Singh, R. P.; Krzystolik, M. G.

2026-08-31 ophthalmology 10.64898/2026.08.29.26361426 medRxiv
Top 1%
0.1%
Show abstract

Purpose: To evaluate whether initiation of GLP-1 receptor agonists (GLP-1RAs) is associated with anti-VEGF treatment burden in type 2 diabetes patients with diabetic macular edema (DME) in the IRIS(R) Registry (Intelligent Research in Sight). Methods: Incident GLP-1RA initiators were matched 1:1 with controls via Mahalanobis distance matching (9,896 pairs; N=19,792) on sociodemographics, DME risk factors, and factors influencing GLP-1RA prescription including hypertension, obesity, chronic kidney disease. A longitudinal mixed-effects event-study model evaluated monthly anti-VEGF injection frequency over a 36-month window (12 months before through 24 months after initiation), adjusting for DME duration. Visual acuity (VA) and central subfield thickness (CST) were secondary outcomes. Results: Following GLP-1RA initiation, anti-VEGF injection trajectories did not significantly differ between the matched GLP-1RA and control cohorts (interaction coefficients -0.18 to 1.59, P>0.05). Likewise, no differences in VA were observed between cohorts (-0.05 to 0.04 logMAR, P>0.05) or CST (-14.12 to 33.58 {micro}m, P>0.05). Conclusion: In these matched cohorts, GLP-1RA initiation was not associated with the trajectory of anti-VEGF use or changes in VA or CST. Precis We used the American Academy of Ophthalmology IRIS(R) Registry (Intelligent Research in Sight) to identify patients with DME. In 19,792 matched patients, there was no significant reduction in injection frequency post GLP1-RA initiation and no significant change in VA or CST.

9
Sex Differences in the Impact of Allosensitization on Waitlist Access and Post-Transplant Outcomes in Adults with Congenital Heart Disease

Joseph, A.; Kearney, K.; Henricks, C.; Morgan, J. L.; Tan, W.; Shafer, K.; Wrobel, C.; Lacelle, C.; Burns, K.; Jawaid, A.; Tapaskar, N.; Solmonson, A.; Nelson, D. B.; Truby, L. K.

2026-09-02 transplantation 10.64898/2026.08.31.26361832 medRxiv
Top 2%
0.0%
Show abstract

Background: Adult congenital heart disease (ACHD) patients are prone to HLA-antibody formation from multiple surgeries, transfusions, and prosthetic surgical material. Females with ACHD may accrue additional, non-surgical alloantigen exposure. Whether sex modifies the impact of allosensitization on heart transplant (HT) access and outcomes in ACHD remains unknown. Methods: We retrospectively analyzed the OPTN/UNOS registry of adults with ACHD listed for first-time HT (2018-2025). Sensitization was defined by calculated panel reactive antibodies (cPRA) at listing. We tested the sex x sensitization (highly sensitized, cPRA >50%) interaction on transplant access using Fine-Gray competing-risks regression, treating transplantation as the event of interest and death or removal from the waitlist as competing events, and on post-transplant survival using multivariable Cox proportional-hazards regression, both adjusted for age at listing, mechanical support at listing, and the number of distinct prior cardiac surgery categories. Results: Among 856 candidates (38% female), females were more often highly sensitized than males (23% vs 14%; age-adjusted OR 1.81, 95% CI 1.26-2.61), even after adjusting for surgical burden. Sensitization reduced transplant access in females (84% to 71%; median wait 60 to 110 days, p < 0.001) but not males (79% vs 79%, median wait 88 vs 98 days). In adjusted Fine-Gray models, the subdistribution hazard for transplant was reduced in sensitized females (sHR 0.54, 95% CI 0.41-0.72) with no effect in males (sHR 0.96, 95% CI 0.73-1.26), and the sex x sensitization interaction was significant (interaction sHR 0.64, 95% CI 0.44-0.94, p = 0.02). Post-transplant mortality was numerically higher in sensitized than non-sensitized candidates in both sexes and the sex x sensitization interaction on 1-year mortality was not significant. The sex-asymmetric effect persisted and was more pronounced in the multiorgan candidates. Conclusions: Allosensitization is not a sex-neutral barrier to transplant in HT candidates with ACHD. Females are more sensitized and have reduced transplant access without differences in 1-year mortality. The female excess in sensitization is not accounted for by surgical burden, and the exposures responsible remain to be defined. These findings warrant a sex-aware listing strategy and further studies.

10
Cross-System Meta-Analysis of Machine Learning Predictors Identifies Value-Specific Risk Drivers and Interactions Underlying Acute Kidney Injury

Chan, H. Y.; Li, D.; Yu, A. S. L.; Kellum, J. A.; Fuhrman, D. Y.; Xu, Q.; Chrischilles, E. A.; Cowell, L. G.; Chandaka, S.; Anzalone, A. J.; Kean, J.; McTigue, K. M.; Mosa, A. S. M.; Taylor, B.; Syed, M.; Waitman, L. R.; Hu, Y.; Liu, M.

2026-09-02 nephrology 10.64898/2026.08.31.26361849 medRxiv
Top 2%
0.0%
Show abstract

Background: Current understanding of acute kidney injury (AKI) risk factors remains largely descriptive, offering limited precision into how specific biomarker values or physiologic thresholds influence susceptibility. We aimed to synthesize knowledge from machine learning models trained across multiple health systems to identify generalizable, value-specific risk drivers and biomarker interactions contributing to AKI risk. Methods: We analyzed electronic health records (EHRs) from 785,497 adult inpatients between 2010 and 2019 across nine U.S. academic medical centers within PCORnet. Interpretable gradient boosting machine models were independently developed at each health system to quantify predictor-outcome associations. Meta-regression was applied to integrate these site-level results, characterize nonlinear value-risk relationships, and identify bivariate interactions between predictors. Results: Meta-analysis revealed consistent, value-specific risk drivers across health systems. An increase in glucose from 100 mg/dL to 140 mg/dL was associated with a 1.46-fold higher risk of AKI. Chloride and anion gap also demonstrated elevated AKI risk with risk increases overlapping portions of their reference ranges, with anion gap showing a 1.14-fold increase across 4-12 mmol/L and chloride a 1.28-fold increase across 96-100 mEq/L. Electrolytes including potassium, calcium, and sodium showed quadratic associations with AKI risk. Bivariate meta-regression identified interactions between key predictors, highlighting pathways that jointly modulate AKI risk. Conclusion: This cross-system meta-analysis synthesizes machine learning-derived evidence into clinically interpretable knowledge, revealing how specific biomarker ranges and interactions modulate AKI risk. By moving beyond surface-level associations to quantitative, generalizable physiologic thresholds, these findings provide actionable insights to enhance risk stratification and personalized prevention in hospital care.

11
Are Frontier Large Language Models Safer Than Government-Backed Symptom Checkers for Clinical Self-Triage? A Standardised Vignette Evaluation

Chowdhury, A. R.; Chowdhury, B.

2026-09-02 health informatics 10.64898/2026.09.01.26361908 medRxiv
Top 2%
0.0%
Show abstract

Background: Consumer use of AI chatbots for health advice is rising, yet triage safety relative to established services remains unclear. Australia's Healthdirect, a government-backed symptom checker with 2.4 million uses in FY2024-25, remains unevaluated against frontier large language models (LLMs), and whether premium subscriptions improve triage safety remains unexplored. This study compared the triage accuracy and safety of Healthdirect against six LLM configurations across ChatGPT, Claude, and Gemini, assessed whether paid subscriptions improve triage safety, and characterised each system's error patterns. Methods: Forty-five clinical vignettes from the Semigran et al. benchmark spanning emergency, non-emergent, and self-care categories (15 each) were evaluated across seven systems. Healthdirect was tested following a seven-rule interaction protocol. LLMs were evaluated using first-person patient-language prompts under free-tier and paid-tier conditions. Outcomes were triage accuracy, emergency sensitivity, under-triage, and critical misses, analysed using Cochran's Q, Bonferroni-corrected McNemar tests, Cohen's kappa, and Wilson intervals. Findings: Triage accuracy differed significantly (Cochran's Q = 36.79, p < 0.001). Healthdirect achieved 48.9% accuracy (95% CI 35.0% to 63.0%; kappa = 0.233) versus 73.3% to 86.7% for LLMs (kappa = 0.600 to 0.800). Healthdirect operated under conservative interactive defaults while LLMs received complete information in a single prompt, which may have disadvantaged Healthdirect. Emergency sensitivity was 46.7% versus 80.0% to 86.7% for LLMs. Healthdirect produced two critical misses; no LLM produced any across 270 evaluations (95% CI 0% to 1.4%). When LLMs undertriaged, they recommended GP care rather than self-care. No tier differences were significant (all p > 0.05), and most systems over-triaged self-care cases. Interpretation: Frontier LLMs demonstrated higher triage accuracy and safer error profiles than Healthdirect. All LLMs avoided critical misses; Healthdirect did not. Premium subscriptions did not significantly improve triage safety. These findings support clinical governance decisions about whether LLMs warrant formal evaluation alongside government-backed symptom checkers.

12
ECG-based longitudinal risk prediction across diseases and organ systems

ye, y.; Zeng, Z.; Tian, X.; Yuan, Z.; Wang, J.; Zhu, Y.

2026-09-02 health informatics 10.64898/2026.08.29.26361697 medRxiv
Top 2%
0.0%
Show abstract

Artificial intelligence applied to routine electrocardiograms (ECGs) has largely focused on detecting existing disease or predicting individual cardiovascular outcomes. Whether ECGs can support prediction of multiple future diseases across organ systems remains unclear. We developed ECG-RISK, a multitask survival model for 67 incident three-character ICD-10 endpoints using ECG waveforms, demographic characteristics and routinely collected laboratory data from 86,673 MIMIC-IV patients. Discrimination was highest for heart, brain, kidney and lung endpoints, with organ-level C-indices ranging from 0.796 to 0.825, whereas liver and pancreatic endpoints showed lower discrimination. The ECG-only model achieved strong discrimination across most endpoints, whereas the incremental improvement gained by incorporating ECG and laboratory inputs beyond demographic information varied substantially across endpoints. Across the nine exploratory aggregated outcomes, Kaplan Meier curves showed clear separation among model-score tertiles. Discrimination was highest for dementia (C-index, 0.891) and heart failure (C-index, 0.857). These findings support the feasibility of ECG-based longitudinal risk prediction across multiple diseases. External validation and competing-risk analyses are required to assess generalisability and clinical utility.

13
Certified large language model-based diagnostic decision support in rheumatology: the ALLIANCE multicentre randomised controlled trial

Kremer, P.; Schlicker, N.; Hasnaj, R.; Bamberger, J.; Witte, T.; Haase, I.; Mayr, A.; Schmidt, C.; Osteras, N.; Baraliakos, X.; Kuhn, S.; Krusche, M.; Knitza, J.

2026-09-02 rheumatology 10.64898/2026.08.29.26361715 medRxiv
Top 2%
0.0%
Show abstract

Objectives To evaluate whether access to a certified large language model (LLM)-based clinical decision support system improves physician diagnostic performance in rheumatology compared with conventional diagnostic resources alone. Methods In this multicentre, open-label, randomised controlled trial, 82 physicians from seven hospitals in two countries were randomised 1:1 to conventional diagnostic resources plus Prof. Valmed or conventional resources alone. Participants assessed three rheumatology vignettes before and after assistance. The primary outcome was top-1 diagnostic accuracy. Secondary outcomes included top-3 accuracy, diagnostic reasoning, confidence, case-processing time and perceived support quality. Results Top-1 accuracy increased from 22.2% to 33.3% in the intervention group and from 23.3% to 35.0% in the control group, with no between-group difference in improvement (adjusted OR 0.99, 95% CI 0.45 to 2.19; p=0.979). Differences in top-3 accuracy, diagnostic reasoning and confidence were also not significant. Assisted case-processing time was substantially shorter with LLM support (94 vs 206 s; adjusted mean difference -112 s, 95% CI -141 to -83; p<0.001). Information timeliness and perceived diagnostic support quality were rated significantly higher in the intervention group. Exploratory analyses showed persistent overconfidence and substantial AI over-reliance. Conclusions Certified LLM-based diagnostic support did not improve diagnostic accuracy compared with conventional resources, but substantially reduced case-processing time and improved perceived support quality. These findings suggest potential workflow benefits while highlighting overconfidence and over-reliance as important safety considerations.

14
Acute Renal, Hepatic, Thromboembolic and Functional Complications after Community-Acquired Acute Lower Respiratory Tract Infection: A Prospective Cohort Study in Bristol, UK, 2022-2024

Chatzilena, A.; Hyams, C.; Challen, R.; Lahuerta, M.; McGuinness, S.; Clout, M.; Begier, E.; King, J.; Morales-Aza, B.; Duale, K.; Rodriguez Pereira, A.; Healy, W.; Southern, J.; Wells, P.; Lihou, K.; Grimes, C.; Campling, J. A.; Maskell, N.; Oliver, J.; Vyse, A.; Gessner, B.; Finn, A.; Danon, L.; The AvonCAP Research Group,

2026-09-02 respiratory medicine 10.64898/2026.08.28.26361617 medRxiv
Top 2%
0.0%
Show abstract

Introduction Acute lower respiratory tract disease (aLRTD) is a leading cause of hospitalisation and death, particularly in older adults and adults with comorbidities, with acute lower respiratory tract infection (aLRTI; pneumonia and non-pneumonic LRTI) being a major component. Non-pulmonary complications and functional decline after aLRTI are recognised, but their pathogen-specific burden is poorly described. We aimed to quantify renal, hepatic, thromboembolic and functional complications, and mortality, after aLRTI hospitalisation, by clinical phenotype and pathogen. Methods We conducted a cohort study of adults (>18 years) admitted with aLRTD to two hospitals in Bristol, UK (01 August 2022-31 July 2024). aLRTD was classified as pneumonia, non-pneumonic LRTI (NP-LRTI) or no diagnosis of aLRTI. Pathogens were identified from standard-of-care and research microbiology. Outcomes were acute kidney injury (AKI), acute liver dysfunction, venous thromboembolism (VTE), in-hospital falls, reduced mobility at discharge, increased care requirements, and 30-day and 1-year mortality. Analyses were descriptive. Results Among 246,797 adult admissions, 21,456 aLRTD hospitalisations were included: 10,239 (47.7%) pneumonia, 7,742 (36.1%) NP-LRTI and 3,475 (16.2%) with no evidence of aLRTI. Of 19,152 tested aLRTD admissions, 8,503 (44.4%) had a positive microbiological/virological test, yielding 9,204 pathogen detections; 1,194 (6.2%) had co-infections, and SARS-CoV-2 was most frequent, with influenza the second most common in pneumonia and NP-LRTI. Pneumonia had greater severity than NP-LRTI and no diagnosis of aLRTI (median length of stay 6 vs 4 vs 4 days; ICU admission 3.4% vs 0.7% vs 0.5%, respectively). Overall, 22.2% developed AKI, 6.1% acute liver dysfunction, 0.6% DVT and 2.4% PE; 1.8% had a fall, 11.5% reduced mobility, and 16.6% required increased care at discharge. 30-day and 1-year mortality were highest for pneumonia (14.0% and 32.0%, respectively). Pathogen-specific analyses showed longer stays and higher complications and mortality rates for SARS-CoV-2 and Streptococcus pneumoniae, and shorter stays with lower complication and mortality rates for influenza and Haemophilus influenzae. Conclusions Non-cardiovascular complications and functional decline after aLRTI were common, particularly in pneumonic and SARS-CoV-2 or pneumococcal disease. These findings support routine surveillance for renal, hepatic, thromboembolic events, early mobilisation and rehabilitation, and consideration of multi-system outcomes when evaluating public health and economic value of vaccines and therapies.

15
Optimizing Aqueous Humor Liquid Biopsy: Safety and Performance of a Short, Low-Dead-Space Ophthalmic Needle for Anterior Chamber Paracentesis

Singh, A. M.; Yeh, T.-C.; DeBoer, C.; Al-Moujahed, A.; Lin, J. B.; Smith, S. J.; Sanislo, S.; Janjua, K. A.; Lin, T.-C.; Almeida, D. R. P.; Mruthyunjaya, P.; Mahajan, V. B.

2026-09-02 ophthalmology 10.64898/2026.08.26.26361364 medRxiv
Top 2%
0.0%
Show abstract

Purpose: To evaluate the safety, procedural performance, sample recovery, and surgeon preference of an ophthalmic needle designed specifically for anterior chamber (AC) paracentesis. Methods: In this multicenter study, AC paracentesis was performed in clinic and operating-room settings using a 32-gauge x 4-mm needle with low dead space. The procedure was evaluated using a standardized physician survey. Prespecified outcomes included procedure-related adverse events (primary outcome), needle entry and handling, aspiration and sample recovery, comparative performance versus a 30-gauge needle, and physician preference for future use. Results: A total of 110 needle uses by eight surgeons were included. No ocular complications occurred, including lens or iris injury, hyphema, AC collapse, wound leak, hypotony, infection, or retinal complication, and no procedure required needle exchange or conversion to another device. Two technical events without ocular sequelae were noted, in which needle entry was partial thickness and did not reach the AC (1.8%; exact 95% CI, 0.2%-6.4%). Physicians rated needle entry, handling and sample recovery as good or excellent. Compared with a 30-gauge needle, the study needle was rated as at least comparable across all assessed domains. All surgeons rated it better or much better for intra-procedural safety and preferred it for future AC taps. Conclusions and Relevance: This short, 32-gauge low-dead-space ophthalmic needle demonstrated a favorable safety profile and was preferred over a 30-gauge needle by all surgeons. As aqueous humor liquid biopsy expands in clinical diagnostics and trials, an ophthalmic-specific needle design may help improve the consistency and safety of aqueous humor collection for molecular analysis and broader clinical use. Keywords: Anterior chamber paracentesis; Aqueous humor; Liquid biopsy; Low dead space; Ophthalmic needle

16
Evaluation of the Efficacy and Safety of Combination Therapy of Vamha and Myrha in the Management of PMOS: An Open-Label, Randomized, Multicentre, Comparative, Prospective Clinical Study

Patil, A.; Barathe, R.; Tate, D. M.; Kate, K.; Pande, S.; Gawande, N.; More, A.; Mahadik, S.; Berde, K.; Singhvi, R.

2026-09-02 sexual and reproductive health 10.64898/2026.08.20.26360875 medRxiv
Top 2%
0.0%
Show abstract

Introduction: Polyendocrine metabolic ovarian syndrome (PMOS), formerly known as polycystic ovary syndrome (PCOS), is a common endocrine disorder affecting women of reproductive age. Besides reproductive and metabolic disturbances, PMOS negatively impacts psychological well-being and quality of life. Despite available treatment options, there remains a need for safe and effective therapies that improve both clinical symptoms and fertility outcomes. Aim: To compare the efficacy of VAMHA and MYRHA tablet combination therapy with standard non-hormonal therapy in restoring regular menstruation. Secondary objectives included assessment of ovulation, menstrual symptoms, polycystic ovarian morphology, hormonal and metabolic parameters, anthropometric measures, and skin manifestations. Study Design: Open-label, randomized, multicentre, prospective comparative clinical study. Methods: Seventy-one women with PMOS were randomized to Group A (n=37) or Group B (n=34). Group A received VAMHA and MYRHA tablets (2 tablets each), while Group B received Metformin 500 mg plus Myoinositol 600 mg (1 tablet), twice daily for 180 days. Data were recorded in Case Report Forms. Statistical Analysis: Continuous variables were summarized using mean and standard deviation, while categorical variables were expressed as frequencies and percentages. Appropriate statistical tests, including Chi-square, were used. A p-value [&le;]0.05 was considered significant. Results: Significantly more participants in Group A achieved regular menstrual cycles than Group B (31 vs. 22; p<0.05). Ovulation occurred in 16 participants in Group A compared with 6 in Group B (p<0.05). Both groups showed significant improvement in menstrual irregularity and related symptoms. Significant reductions in Anti-Mullerian Hormone (AMH), fasting insulin, and body mass index (BMI) were observed in both groups (p<0.05). Resolution of polycystic ovarian morphology occurred in 13 participants (38.23%) in Group A and 10 (33.33%) in Group B. Both treatments were well tolerated with no major safety concerns. Conclusions: VAMHA and MYRHA combination therapy was superior to standard non-hormonal therapy in improving menstrual regularity and ovulation. It also produced favourable metabolic, hormonal, and ultrasonographic outcomes, suggesting its potential as a safe and effective option for comprehensive PMOS management and fertility enhancement.

17
Burden of fatigue in compensated chronic liver disease: findings from the multinational a:GAP Study

Choudhuri, G.; Akhundova-Unadkat, G.; Naidoo, N.; Morales-Castillo, M.; Guillaume, X.; Duijnhoven, R. G.; Safaei, A.; Swain, M. G.

2026-09-02 gastroenterology 10.64898/2026.08.28.26361618 medRxiv
Top 2%
0.0%
Show abstract

Background & Aims: Fatigue is a central symptom of chronic liver disease (CLD), substantially impacting health-related quality of life (HRQoL). This study aimed to further understand CLD symptomatology, including fatigue, and its impact on HRQoL from a patient perspective. Methods: Abbott Global Assessment of Patients unmet needs (aGAP) was a multinational, cross-sectional survey in adults with compensated CLD in China, India and Mexico, conducted between July and November 2024. Adult participants who self-reported that they had physician-diagnosed CLD and were experiencing fatigue completed a quantitative survey to assess symptom burden and included three HRQoL patient-reported outcome (PRO) questionnaires (Patient-Reported Outcomes Measurement Information System [PROMIS]-29+2, Work Productivity and Activity Impairment - Specific Health Problem version 2.0 [WPAI: SHP], Multidimensional Fatigue Inventory [MFI]). Results: Overall, 505 participants (China: 200; Mexico: 105; India: 200) completed the study. Participants reported that their CLD-related fatigue sometimes, often or always affected their self-esteem/confidence (45.1%) and ability to maintain or acquire new employment (38.6%). Most participants reported moderate (51.3%) or serious (26.9%) fatigue, with 33.5% experiencing fatigue every day or almost every day. Many participants felt their social life was negatively impacted by their fatigue (47.3%) and that there were related financial difficulties (53.9%). Use of validated PRO tools demonstrated severe fatigue (MFI: overall mean [SD] 13.9 [3.4] general fatigue and 13.4 [3.6] physical fatigue) as well as substantial levels of work and activity impairment (WPAI: SHP overall mean [SD] 53.0 [26.4]) and high levels of anxiety, pain interference, depression and sleep interference (PROMIS T-scores [&ge;]54). Conclusions: Fatigue has a substantial impact on HRQoL among adults with CLD across several countries, highlighting a global unmet need for targeted interventions to effectively identify and manage the condition.

18
GLP-1/GIP Uptake, Indication, and Access Pathways Among US Adults in the Understanding America Study

Chaturvedi, R. R.; Gracner, T.; Perez-Arce, F.; Suen, S.-c.; Jin, J.; Orriens, B.; Pacula, R. L.; Sexton Ward, A.; Haile, R.; Kapteyn, A.

2026-09-02 endocrinology 10.64898/2026.08.28.26361368 medRxiv
Top 2%
0.0%
Show abstract

Importance: Evidence on GLP-1/GIP therapies is largely derived from trials enrolling selected populations or medical records that miss utilization outside healthcare channels. No nationally representative cohort has characterized real-world uptake, indications, and access. Objective: To characterize GLP-1/GIP prevalence, indication, clinical profile, and access. Design: Prospective cohort study with three GLP-1/GIP surveillance waves (March 2024, December 2024, October 2025). Setting: The Understanding America Study, an address-based, nationally representative panel of approximately 15,000 US adults aged 18+ years initiated in 2014. Participants: UAS participants responding to at least one surveillance wave (n=9150). Exposures: GLP-1/GIP use status (never vs any use, comprising current and former use), self-reported primary indication (diabetes, weight loss, or other), and access pathway (traditional vs non-traditional). Main Outcomes and Measures: Survey-weighted prevalence of GLP-1/GIP use, overall and by indication and access pathway; sociodemographic, cardiometabolic, treatment, and access characteristics; and smartwatch-derived resting heart rate, heart rate variability, maximum activity heart rate, step count, and sleep duration and variability. Results: Among n=9150 adults (1274 with any use; 60.9% female; median age 53 years), weighted prevalence increased 46%, from 8.2% (March 2024) to 12.0% (October 2025) representing 32 million. Weight-loss indications grew, reaching nearly half of use (4.1% to 5.6%); diabetes-indicated use was stable (5.3% to 5.4%). Users carried high cardiometabolic burden (obesity, 68.2%; diabetes, 53.6%) but diverged by indication: diabetes-indicated users were older (median, 59 vs 49 years), whereas weight-loss-indicated users were more often female (69.9% vs 51.3%) and healthier. One in three users (~9 million) had non-traditional access, especially in weight-loss-indicated users, of whom 33% had no conventional prescription; 41% used compounding, online, or foreign pharmacies; and, 43% lacked coverage. Non-traditional users were five times as likely to report an unlisted, likely compounded formulation (19.8% vs 4.1%). All p<0.05. Conclusions and Relevance: Real-world GLP-1/GIP use has grown rapidly and diversified substantially in indication, access, and population profile. One in 3 users obtained treatment through nontraditional channels largely invisible to claims data, raising long-term safety, efficacy, and coverage questions. GLIMMER provides a public, nationally representative longitudinal evidence base for future payer and provider decisions.

19
An interpretable, formally verified point-of-care ultrasound risk equation for difficult videolaryngoscopy: development and internal validation

Oyarzun-Silva, R. A.; Hernandez-Hernandez, P.; Fernandez-Vaquero, M. A.; De Luis-Cabezon, N.

2026-09-02 anesthesia 10.64898/2026.08.28.26361621 medRxiv
Top 2%
0.0%
Show abstract

Background. Videolaryngoscopy still requires adjuncts or hyperangulated rescue in a clinically important minority, and bedside screening discriminates modestly. Point-of-care ultrasound (POCUS) of the anterior airway is a promising alternative, but existing prediction models are opaque or assume a pre-specified functional form. We developed and internally validated a parsimonious, fully disclosed POCUS risk equation whose form is recovered from data and whose structural properties are machine-checked by formal proof - to our knowledge the first formally verified clinical risk predictor - following TRIPOD+AI 2024. Methods. In a prospective single-centre, single-operator cohort of 259 adults undergoing elective videolaryngoscopy (no-Easy airway 68/259, 26.3%), Sequentially Thresholded Least Squares with bootstrap stability selection (B=300) screened a 71-term library of nine POCUS features and retained a seven-term logistic equation; a two-term bootstrap-stable model was pre-specified as robustness analysis. Internal validation used 5x10 repeated cross-validation plus temporal and device hold-outs, with pre-specified overfitting and optimism assessments. Five behavioural properties of the deployed equation were machine-checked in Lean 4. Results. Two interactions met the |c|/sigma_c>2 stability criterion: skin-to-epiglottis x skin-to-hyoid-bone distance and tongue volume x sagittal tongue area. The seven-term equation reached a 5x10 cross-validated C-statistic of 0.966 (optimism-corrected 0.968) and held across temporal and device hold-outs (0.94-0.97). Calibration-in-the-large matched prevalence, with cross-validated slope 0.90 attenuating to 0.625 out-of-time; standard recalibration restored 0.92 without loss of discrimination. The pre-specified two-term robustness model reproduced this performance (C-statistic 0.964-0.968; events-per-parameter 34; shrinkage 0.99), confirming the result is not an artefact of the screening stage. Net benefit over a clinical baseline was positive across 10-50% thresholds. All five Lean 4 theorems compiled without sorry. Conclusions. A sparse, formally verified POCUS equation predicts difficult videolaryngoscopy with high internally validated discrimination and quantified, modest overfitting. Because the equation was developed in a single-operator cohort and its inputs are operator-dependent, external validation requires prior harmonisation of the measurement protocol and operator credentialing.

20
Lung function trajectories in children with cystic fibrosis aged 3-17 years: impact of elexacaftor-tezacaftor-ivacaftor on lung function

Dyer, B. P.; Deery, M.; Heyman, R.; Robinson, P.; Wainwright, C.; Sly, P.; Ware, R.; Blake, T.

2026-09-02 respiratory medicine 10.64898/2026.08.31.26361791 medRxiv
Top 2%
0.0%
Show abstract

Background Elexacaftor-tezacaftor-ivacaftor (ETI) has been demonstrated to improve lung function in clinical trials; however, evidence describing effects on trajectories and whether long-term improvements are sustained (>1-year) is lacking. We estimated within-person lung clearance index (LCI) trajectories before and after ETI initiation, assessing changes in level and rate of change, alongside acute LCI change, up to three years after ETI initiation. Methods Prospective observational study of children at a tertiary hospital. Children aged 3-17 years with [&ge;]2 LCI testing occasions (i) before and (ii) after starting ETI were used to describe lung function trajectories. Children with [&ge;]1 pre-ETI and [&ge;]1 post-ETI LCI occasion(s) were used to describe acute LCI change after ETI initiation. Age-adjusted LCI trajectories for time periods (i) before and (ii) after ETI initiation were estimated using linear mixed-effects models, and pre- and post-ETI LCIs were compared using paired Wilcoxon tests. Results Mean pre-ETI and post-ETI longitudinal changes in LCI were -0.007 (95% CI: -0.28, 0.27; n=35) and 0.12 (95% CI: -0.17, 0.41; n=20) turnovers per year, respectively. Before ETI initiation, 57% (30/53) of patients had an LCI[&ge;]7.1 turnovers (indicating impaired lung function), compared to 26% (14/53) post-ETI, with a median LCI difference of -0.70 (95% CI -0.84, -0.46; p<0.001) turnovers. Within-individual variability in LCI decreased post-ETI. Conclusions Our real-world data within a unique longitudinal study provide a comprehensive picture of ETI benefit by outlining not only acute improvement in LCI but maintained stability in LCI trajectories and improved LCI stability sustained up to three years post-initiation.